Back

Journal of Pathology Informatics

Elsevier BV

Preprints posted in the last 7 days, ranked by how well they match Journal of Pathology Informatics's content profile, based on 15 papers previously published here. The average preprint has a 0.02% match score for this journal, so anything above that is already an above-average fit.

1
BanffNET, a Deep Learning System for Comprehensive Histological Lesion Quantification in Kidney Transplant Biopsies

Buzzanca, G.; Pala, C.; He, J.; Hofstraat-Boersma, R.; Tammaro, A.; van Midden, D.; Buelow, R.; Hoelscher, D. L.; Muehlfeld, A. S.; Koeller, m.; Kozakowski, N.; Boehmig, G.; Halloran, P. F.; van der Helm, D.; Meziyerh, S.; Venhuizen, J.-H.; Haitjema, S.; Dijkstra, J.; Hilbrands, L. B.; Steenbergen, E. J.; van Zuilen, A. D.; Nurmohamed, A. S.; Bemelman, F. J.; Bruns, I. B.; Callegaro, G.; van de Water, B.; Pieters, T. T.; Breimer, G. E.; Rossi, G. M.; Fiaccadori, E.; Maggiore, U.; Roelofs, J. J. T. H.; Testa, F.; Fontana, F.; Abiola, A. A.; Delsante, M.; Corthals, G. L.; Peters-Sengers, H.; Ngu

2026-09-02 pathology 10.64898/2026.08.28.26360029 medRxiv
Top 0.2%
1.1%
Show abstract

Accurate, reproducible interpretation of kidney allograft biopsies is critical for diagnosis of graft injury to guide prognosis and management. The international Banff classification is a consensus diagnostic system based on semiquantitative histological lesion scoring on either extent or severity of kidney transplant biopsies. However, pathologist scoring is limited by substantial interobserver variability, constrained scalability, and the inherent nature of the scoring system itself. Here we present BanffNET, a weakly supervised, probabilistic deep learning framework that combines self-supervised feature extraction with a novel Bayesian multiple-instance learning framework to predict (continuously) the full spectrum of Banff lesion scores directly from whole-slide images (WSIs). Using lesion-specific aggregation functions tailored to localized (modeling lesion severity) and diffuse pathologies (modeling lesion extent), BanffNET generates interpretable, patch-level probability maps and calibrated slide-level scores. BanffNET's performance was assessed relative to consensus, biological correlates of rejection and clinical outcome, demonstrating superior consistency, transportability and generalization. Trained on 7,249 WSIs from three cohorts, BanffNET demonstrates consistent performance on 11,028 WSIs across five external test sets, performing on par or exceeding expert consensus across lesions. BanffNET scores align more closely than pathologist Banff scores with molecular profiles of rejection, offering a transparent, biologically grounded framework for computational pathology with relevance beyond transplantation.

2
Plasma and follicular fluid concentrations of carotenoids, tocopherols and retinol in a French population of women undergoing in vitro fertilization: a monocentric non-interventional study

Ndiaye, A.; Thiebaut, A. C. M.; Borel, P.; Sabran, C.; Elis, S.; Guerif, F.; Maillard, V.

2026-09-01 sexual and reproductive health 10.64898/2026.08.28.26360803 medRxiv
Top 0.2%
1.0%
Show abstract

The distribution of fat-soluble compounds (including antioxidants) in follicular fluid (FF) remains sparsely documented in relation to in vitro fertilization (IVF) outcomes and existing studies have reported diverging associations. This study aimed to describe plasma and FF concentrations of fat-soluble micronutrients in women undergoing IVF and to analyze their adjusted associations with ovarian function, embryo development and pregnancy outcomes. In 2021-2022, plasma and FF samples were collected from 82 women (first IVF cycle) at oocyte puncture, along with lifestyle data covering the three preceding months. Eleven compounds (two tocopherols, three xanthophylls, five carotenes and retinol) were quantified. All compounds were detected in both compartments (lowest in FF) except phytoene, undetectable in FF. Plasma and FF -tocopherol concentrations were positively associated with plasma estradiol levels before oocyte puncture (both p<0.01) while FF -carotene and lycopene were inversely associated with plasma progesterone concentrations (p=0.01 and 0.02, respectively). Plasma phytofluene and phytoene were positively associated with mature oocyte rate (p=0.03 and p=0.01, respectively), while FF retinol was negatively associated (p=0.03). Carotenes, tocopherols and retinol were inversely associated with later IVF outcomes: fertilization rate (p<0.001 for plasma g-tocopherol, 0.02 for FF retinol), top-quality embryo (p=0.02 for plasma phytofluene), biochemical pregnancy at day 7 post-embryo transfer (p=0.05 for plasma -tocopherol, 0.02 for plasma -carotene), clinical pregnancy (p=0.03 for plasma -tocopherol, 0.01 for plasma phytoene) and live birth (p=0.04 for plasma -tocopherol, 0.02 for plasma phytoene). Plasma and FF g-tocopherol were positively associated with embryo fragmentation (both p<0.05). Finally, among xanthophylls, only plasma {beta}-cryptoxanthin was positively associated with plasma progesterone concentrations (p=0.02). Our findings of heterogeneous associations between tocopherols, carotenes, retinol and IVF outcomes across the stages of IVF suggest a beneficial effect limited to early outcomes and support a complex and context-dependent role of these compounds in female reproduction. This manuscript has been submitted to PlosOne on August 19, 2026.

3
AI-Driven Early Detection of Polycystic Ovary Syndrome via Follicle Count

Thota, D.; Mahesha, A.; Khasim, M. F.; Kethineni, K. P.; Pothireddygari, B.; Rahmani, B.

2026-09-04 health informatics 10.64898/2026.09.01.26361974 medRxiv
Top 0.3%
0.6%
Show abstract

Polycystic Ovary Syndrome is a common endocrine disorder characterized by ovulatory dysfunction, hyperandrogenism, and/or polycystic ovarian morphology, with significant reproductive and metabolic consequences. Due to heterogeneous symptom profiles, Polycystic Ovary Syndrome is frequently underdiagnosed or diagnosed late. In this study, we develop machine learning models for early Polycystic Ovary Syndrome prediction using a structured clinical dataset with 42 features and 542 patient records. After data cleaning and normalization, correlation-based feature selection was applied to retain the most predictive variables. Multiple models were trained and evaluated, including Logistic Regression, Decision Tree, KNN, and Random Forest. Results demonstrate that Random Forest achieves the best overall performance (approximately 88% accuracy), suggesting that ensemble models can effectively capture non-linear feature interactions in clinical data. We also contextualize findings with international clinical guidance and recent work on explainable and clinically applicable Polycystic Ovary Syndrome prediction systems.

4
Hybrid risk scores integrating polygenic and clinical variables for endometriosis prediction

Goroshchuk, O.; Koller, D.

2026-09-03 epidemiology 10.64898/2026.08.31.26361798 medRxiv
Top 0.4%
0.5%
Show abstract

Background: Endometriosis affects approximately 10% of reproductive-age women and is associated with substantial diagnostic delay and heterogeneous symptom presentation. Prior machine-learning prediction models have relied on comorbidity data alone or on small candidate-variant genetic scores, with inconsistent or incompletely reported performance. No study has combined a well-powered, multi-ancestry polygenic risk score (PRS) with environmental, reproductive, and symptom data in a single hybrid model. We developed and evaluated hybrid risk-prediction models integrating a genome-wide, multi-ancestry PRS with clinical and symptom data for endometriosis in the US-based All of Us Research Program. Methods: Among 69,376 participants (15,382 endometriosis cases, 53,994 controls) across six genetically inferred ancestry groups, we computed individual-level PRS values using PRS-CS weights derived from an independent, multi-ancestry GWAS. Five nested logistic regression, random forest, and XGBoost models progressively added age, ancestry, and within-ancestry genetic principal components (Model 1), environmental and reproductive factors (Model 2), symptom and comorbidity indicators (Model 3), all covariates combined (Model 4), and PRS x environment interactions (Model 5). Performance was assessed by AUROC in a held-out test set and 5-fold cross-validation, with class-weighted, Youden-optimized thresholds used for sensitivity, specificity, and predictive values; permutation importance identified top contributors. Pairwise AUROC differences were tested with a Holm-corrected DeLong-type test. Results: Discrimination improved from AUROC 0.63 (PRS, age, ancestry, principal components) to 0.72 for the full model, driven mainly by symptom and comorbidity data. XGBoost consistently outperformed logistic regression and random forest. The PRS ranked among the top individual predictors by permutation importance in nearly every model, alongside age, while genetic and demographic information alone gave only modest discrimination, and PRS x environment interactions did not improve on environmental factors alone. Threshold optimization yielded balanced sensitivity and specificity (~0.67/0.65) versus near-zero sensitivity at a default threshold. Conclusions: Combining the PRS with symptom and comorbidity data gave the best discrimination compared to solely a well-powered, multi-ancestry PRS as a predictor of endometriosis. This study clarifies both the promise and current limits of hybrid genetic-clinical prediction for endometriosis and points to symptom-based phenotyping, molecular subtyping, and external validation as priorities.

5
Collagen staining with fast green FCF enables 3D imaging of pulmonary fibrosis

Saqib, M.; Rivers, A. K.; Masala, S.; Baker, J. R.; Hobbs, C.; Boden, A.; Jose, A. A.; Herzog, D.; Cleary, S. J.

2026-08-31 pathology 10.64898/2026.08.27.747478 medRxiv
Top 0.4%
0.4%
Show abstract

Current approaches for imaging fibrotic remodeling have sensitivity, specificity and cost drawbacks that limit both preclinical research and clinical diagnosis. Here, we show that fast green FCF, a small molecule that binds to fibrillar collagen, enables highly sensitive and specific imaging of fibrosis in lung samples from mice and humans using fluorescence microscopy. We report strategies for using fast green FCF staining to assess fibrotic remodeling using precision-cut lung slice and whole-biopsy preparations. Our findings demonstrate that fluorescence imaging of fast green FCF-stained collagen will be useful for fibrosis research and may help to improve detection of fibrosis in clinical pathology.

6
Mural-VISTA: A tool for mural cell-vessel interaction assessment and multiscale single-cell topo-morphological analysis

Zeng, H.; Hu, M.; Phng, L.-K.; Matsunaga, Y. T.

2026-09-01 bioinformatics 10.64898/2026.08.27.747487 medRxiv
Top 0.6%
0.1%
Show abstract

Three-dimensional (3D) mural cell morphology is heterogeneous and coupled to vessel geometry, however, measurements from two-dimensional (2D) maximum intensity projections (MIP) obscure overlapping processes and cell-vessel contacts. Accordingly, we developed Mural-VISTA, a semi-automated Python workflow for mural cell-vessel interaction and single-cell topo-morphology analysis of reconstructed surface meshes. This workflow integrates mesh pretreatment, interactive centerline extraction, hierarchical segmentation of cell soma, main axis and secondary processes (branches), and extraction of 36 multiscale (cell process segment level, process level, and whole cell level) topo-morphological and vessel-referenced metrics. Mural-VISTA identified morphological changes in pericytes and vascular smooth muscle cells (vSMCs) with altered RhoA activity. Constitutive active RhoA (RhoA CA) over-expression reduced branch complexity and increased process alignment in both cell types, while increased whole-cell and branch solidity only in vSMCs. Dominant negative RhoA (RhoA DN) over-expression increased branch abundance and reduced branch solidity in pericytes but not vSMCs, suggesting cell-type specific effect of reduced RhoA activity. In conclusion, Mural-VISTA enables quantitative 3D profiling of mural cell architecture and its spatial relationship with the vessel.

7
Skin Cancer Classification Using Explainable Artificial Intelligence With an Ensemble Model and Rigorous Leakage Free Validation

BARAN, M. T.; KARAKOYUN, O.

2026-09-05 health systems and quality improvement 10.64898/2026.09.02.26362011 medRxiv
Top 0.7%
0.1%
Show abstract

Background: Reliable melanoma classification requires models that capture both local dermoscopic morphology and broader contextual patterns while maintaining auditable, leakageaware internal validation. Objectives: To develop and internally validate an EfficientNetB0-Swin Transformer Tiny ensemble for classifying histopathologically verified dermoscopic images as benign melanocytic lesions or malignant melanoma. Methods: This retrospective diagnostic model-development and internal validation study screened 552,869 ISIC Archive records; filtering and dermatologist review yielded 1,199 uniquepatient and unique lesion images (578 benign and 621 malignant). Images were the predictors and histopathology was the reference. ImageNet pretrained EfficientNetB0 and Swin-T features were fused. Patient independent five fold validation used weighted sampling, mixup, label smoothing, AdamW, early stopping, and five view test time augmentation. Results: Mean accuracy was 0.89325 {+/-} 0.03179, mean receiver operating characteristic area under the curve (ROC-AUC) was 0.96348 {+/-} 0.01695, and mean support weighted F1-score was 0.89300 {+/-} 0.03220. The fold level 95% confidence intervals were 0.8538-0.9327 for accuracy and 0.9424-0.9845 for ROC-AUC. Pooled counts were 526 true negatives, 52 false positives, 76 false negatives, and 545 true positives, yielding 87.76% sensitivity and 91.00% specificity. Qualitative Grad-CAM review showed peripheral artifact activation in two false positives and lesion centered activation in two correctly classified cases; these observations were not systematically scored. Limitations: The validation folds were also used for early stopping and checkpoint selection. Device stratified analysis, systematic interpretability scoring, calibration, and independent external validation were unavailable. Conclusions: The ensemble showed high internal discrimination and is intended only as a clinician facing adjunct. The error audit workflow enables targeted retrospective review, but external validation is required before clinical use or generalizability claims.

8
GLP-1/GIP Uptake, Indication, and Access Pathways Among US Adults in the Understanding America Study

Chaturvedi, R. R.; Gracner, T.; Perez-Arce, F.; Suen, S.-c.; Jin, J.; Orriens, B.; Pacula, R. L.; Sexton Ward, A.; Haile, R.; Kapteyn, A.

2026-09-02 endocrinology 10.64898/2026.08.28.26361368 medRxiv
Top 1%
0.0%
Show abstract

Importance: Evidence on GLP-1/GIP therapies is largely derived from trials enrolling selected populations or medical records that miss utilization outside healthcare channels. No nationally representative cohort has characterized real-world uptake, indications, and access. Objective: To characterize GLP-1/GIP prevalence, indication, clinical profile, and access. Design: Prospective cohort study with three GLP-1/GIP surveillance waves (March 2024, December 2024, October 2025). Setting: The Understanding America Study, an address-based, nationally representative panel of approximately 15,000 US adults aged 18+ years initiated in 2014. Participants: UAS participants responding to at least one surveillance wave (n=9150). Exposures: GLP-1/GIP use status (never vs any use, comprising current and former use), self-reported primary indication (diabetes, weight loss, or other), and access pathway (traditional vs non-traditional). Main Outcomes and Measures: Survey-weighted prevalence of GLP-1/GIP use, overall and by indication and access pathway; sociodemographic, cardiometabolic, treatment, and access characteristics; and smartwatch-derived resting heart rate, heart rate variability, maximum activity heart rate, step count, and sleep duration and variability. Results: Among n=9150 adults (1274 with any use; 60.9% female; median age 53 years), weighted prevalence increased 46%, from 8.2% (March 2024) to 12.0% (October 2025) representing 32 million. Weight-loss indications grew, reaching nearly half of use (4.1% to 5.6%); diabetes-indicated use was stable (5.3% to 5.4%). Users carried high cardiometabolic burden (obesity, 68.2%; diabetes, 53.6%) but diverged by indication: diabetes-indicated users were older (median, 59 vs 49 years), whereas weight-loss-indicated users were more often female (69.9% vs 51.3%) and healthier. One in three users (~9 million) had non-traditional access, especially in weight-loss-indicated users, of whom 33% had no conventional prescription; 41% used compounding, online, or foreign pharmacies; and, 43% lacked coverage. Non-traditional users were five times as likely to report an unlisted, likely compounded formulation (19.8% vs 4.1%). All p<0.05. Conclusions and Relevance: Real-world GLP-1/GIP use has grown rapidly and diversified substantially in indication, access, and population profile. One in 3 users obtained treatment through nontraditional channels largely invisible to claims data, raising long-term safety, efficacy, and coverage questions. GLIMMER provides a public, nationally representative longitudinal evidence base for future payer and provider decisions.

9
Evaluating Mean Platelet Volume in relation to Disease Severity in Paediatric Sickle Cell Anaemia: A Cross-Sectional Study in Kwara State, North-Central Nigeria

Oladimeji, F. D.; Adewoyin, A. D.; Oyeleke, K. O.

2026-09-02 hematology 10.64898/2026.08.28.26361349 medRxiv
Top 1%
0.0%
Show abstract

Background: Sickle cell anaemia (SCA) is characterised by chronic haemolysis, inflammation, platelet activation, and recurrent vaso-occlusive complications. Mean platelet volume (MPV) is a readily available platelet index, but evidence regarding its relationship with disease severity in paediatric SCA remains limited and inconsistent, particularly in African populations. Objective: To evaluate the relationship between MPV and disease severity among children with SCA in Kwara State, North-Central Nigeria. Methods: This hospital-based cross-sectional study included 51 clinically stable children with confirmed SCA consecutively recruited from the paediatric haematology clinic of Children Emergency Specialist Hospital, Ilorin. Complete blood count, including MPV, was performed using a Rayto RT-7600 automated haematology analyser. Disease severity was assessed using a composite clinical and laboratory scoring system based on a previously described method. Pearson's correlation, Spearman's rank correlation, simple linear regression, and the Kruskal-Wallis test were used as appropriate. Statistical significance was set at p < 0.05. Results: Of 51 participants, 14 (27.5%) had mild, 33 (64.7%) moderate, and 4 (7.8%) severe disease. Mean MPV was 9.34 +/- 0.76 fL (range, 8.0-11.2). Pearson's correlation showed a weak positive, non-significant linear relationship with severity score (r = 0.231, p = 0.103), whereas Spearman's analysis showed a weak positive monotonic association (rho = 0.286, p = 0.042). Regression explained 5.3% of severity-score variation (R2 = 0.053, p = 0.103). MPV did not differ significantly across severity categories (H = 2.163, p = 0.339). MPV correlated inversely with haemoglobin (r = -0.556, p < 0.001) and positively with platelet count (r = 0.307, p = 0.029). Conclusion: MPV showed a weak relationship with disease severity but inconsistent statistical evidence across analyses. The limited explained variance and absence of significant differences between severity categories do not support MPV as a standalone severity marker. Larger longitudinal studies are warranted. Keywords: Sickle cell anaemia; Mean platelet volume; Disease severity; Platelet indices; Paediatric haematology; Cross-sectional study; Nigeria.

10
Are Frontier Large Language Models Safer Than Government-Backed Symptom Checkers for Clinical Self-Triage? A Standardised Vignette Evaluation

Chowdhury, A. R.; Chowdhury, B.

2026-09-02 health informatics 10.64898/2026.09.01.26361908 medRxiv
Top 1%
0.0%
Show abstract

Background: Consumer use of AI chatbots for health advice is rising, yet triage safety relative to established services remains unclear. Australia's Healthdirect, a government-backed symptom checker with 2.4 million uses in FY2024-25, remains unevaluated against frontier large language models (LLMs), and whether premium subscriptions improve triage safety remains unexplored. This study compared the triage accuracy and safety of Healthdirect against six LLM configurations across ChatGPT, Claude, and Gemini, assessed whether paid subscriptions improve triage safety, and characterised each system's error patterns. Methods: Forty-five clinical vignettes from the Semigran et al. benchmark spanning emergency, non-emergent, and self-care categories (15 each) were evaluated across seven systems. Healthdirect was tested following a seven-rule interaction protocol. LLMs were evaluated using first-person patient-language prompts under free-tier and paid-tier conditions. Outcomes were triage accuracy, emergency sensitivity, under-triage, and critical misses, analysed using Cochran's Q, Bonferroni-corrected McNemar tests, Cohen's kappa, and Wilson intervals. Findings: Triage accuracy differed significantly (Cochran's Q = 36.79, p < 0.001). Healthdirect achieved 48.9% accuracy (95% CI 35.0% to 63.0%; kappa = 0.233) versus 73.3% to 86.7% for LLMs (kappa = 0.600 to 0.800). Healthdirect operated under conservative interactive defaults while LLMs received complete information in a single prompt, which may have disadvantaged Healthdirect. Emergency sensitivity was 46.7% versus 80.0% to 86.7% for LLMs. Healthdirect produced two critical misses; no LLM produced any across 270 evaluations (95% CI 0% to 1.4%). When LLMs undertriaged, they recommended GP care rather than self-care. No tier differences were significant (all p > 0.05), and most systems over-triaged self-care cases. Interpretation: Frontier LLMs demonstrated higher triage accuracy and safer error profiles than Healthdirect. All LLMs avoided critical misses; Healthdirect did not. Premium subscriptions did not significantly improve triage safety. These findings support clinical governance decisions about whether LLMs warrant formal evaluation alongside government-backed symptom checkers.

11
ECG-based longitudinal risk prediction across diseases and organ systems

ye, y.; Zeng, Z.; Tian, X.; Yuan, Z.; Wang, J.; Zhu, Y.

2026-09-02 health informatics 10.64898/2026.08.29.26361697 medRxiv
Top 1%
0.0%
Show abstract

Artificial intelligence applied to routine electrocardiograms (ECGs) has largely focused on detecting existing disease or predicting individual cardiovascular outcomes. Whether ECGs can support prediction of multiple future diseases across organ systems remains unclear. We developed ECG-RISK, a multitask survival model for 67 incident three-character ICD-10 endpoints using ECG waveforms, demographic characteristics and routinely collected laboratory data from 86,673 MIMIC-IV patients. Discrimination was highest for heart, brain, kidney and lung endpoints, with organ-level C-indices ranging from 0.796 to 0.825, whereas liver and pancreatic endpoints showed lower discrimination. The ECG-only model achieved strong discrimination across most endpoints, whereas the incremental improvement gained by incorporating ECG and laboratory inputs beyond demographic information varied substantially across endpoints. Across the nine exploratory aggregated outcomes, Kaplan Meier curves showed clear separation among model-score tertiles. Discrimination was highest for dementia (C-index, 0.891) and heart failure (C-index, 0.857). These findings support the feasibility of ECG-based longitudinal risk prediction across multiple diseases. External validation and competing-risk analyses are required to assess generalisability and clinical utility.

12
Certified large language model-based diagnostic decision support in rheumatology: the ALLIANCE multicentre randomised controlled trial

Kremer, P.; Schlicker, N.; Hasnaj, R.; Bamberger, J.; Witte, T.; Haase, I.; Mayr, A.; Schmidt, C.; Osteras, N.; Baraliakos, X.; Kuhn, S.; Krusche, M.; Knitza, J.

2026-09-02 rheumatology 10.64898/2026.08.29.26361715 medRxiv
Top 1%
0.0%
Show abstract

Objectives To evaluate whether access to a certified large language model (LLM)-based clinical decision support system improves physician diagnostic performance in rheumatology compared with conventional diagnostic resources alone. Methods In this multicentre, open-label, randomised controlled trial, 82 physicians from seven hospitals in two countries were randomised 1:1 to conventional diagnostic resources plus Prof. Valmed or conventional resources alone. Participants assessed three rheumatology vignettes before and after assistance. The primary outcome was top-1 diagnostic accuracy. Secondary outcomes included top-3 accuracy, diagnostic reasoning, confidence, case-processing time and perceived support quality. Results Top-1 accuracy increased from 22.2% to 33.3% in the intervention group and from 23.3% to 35.0% in the control group, with no between-group difference in improvement (adjusted OR 0.99, 95% CI 0.45 to 2.19; p=0.979). Differences in top-3 accuracy, diagnostic reasoning and confidence were also not significant. Assisted case-processing time was substantially shorter with LLM support (94 vs 206 s; adjusted mean difference -112 s, 95% CI -141 to -83; p<0.001). Information timeliness and perceived diagnostic support quality were rated significantly higher in the intervention group. Exploratory analyses showed persistent overconfidence and substantial AI over-reliance. Conclusions Certified LLM-based diagnostic support did not improve diagnostic accuracy compared with conventional resources, but substantially reduced case-processing time and improved perceived support quality. These findings suggest potential workflow benefits while highlighting overconfidence and over-reliance as important safety considerations.

13
Optimizing Aqueous Humor Liquid Biopsy: Safety and Performance of a Short, Low-Dead-Space Ophthalmic Needle for Anterior Chamber Paracentesis

Singh, A. M.; Yeh, T.-C.; DeBoer, C.; Al-Moujahed, A.; Lin, J. B.; Smith, S. J.; Sanislo, S.; Janjua, K. A.; Lin, T.-C.; Almeida, D. R. P.; Mruthyunjaya, P.; Mahajan, V. B.

2026-09-02 ophthalmology 10.64898/2026.08.26.26361364 medRxiv
Top 1%
0.0%
Show abstract

Purpose: To evaluate the safety, procedural performance, sample recovery, and surgeon preference of an ophthalmic needle designed specifically for anterior chamber (AC) paracentesis. Methods: In this multicenter study, AC paracentesis was performed in clinic and operating-room settings using a 32-gauge x 4-mm needle with low dead space. The procedure was evaluated using a standardized physician survey. Prespecified outcomes included procedure-related adverse events (primary outcome), needle entry and handling, aspiration and sample recovery, comparative performance versus a 30-gauge needle, and physician preference for future use. Results: A total of 110 needle uses by eight surgeons were included. No ocular complications occurred, including lens or iris injury, hyphema, AC collapse, wound leak, hypotony, infection, or retinal complication, and no procedure required needle exchange or conversion to another device. Two technical events without ocular sequelae were noted, in which needle entry was partial thickness and did not reach the AC (1.8%; exact 95% CI, 0.2%-6.4%). Physicians rated needle entry, handling and sample recovery as good or excellent. Compared with a 30-gauge needle, the study needle was rated as at least comparable across all assessed domains. All surgeons rated it better or much better for intra-procedural safety and preferred it for future AC taps. Conclusions and Relevance: This short, 32-gauge low-dead-space ophthalmic needle demonstrated a favorable safety profile and was preferred over a 30-gauge needle by all surgeons. As aqueous humor liquid biopsy expands in clinical diagnostics and trials, an ophthalmic-specific needle design may help improve the consistency and safety of aqueous humor collection for molecular analysis and broader clinical use. Keywords: Anterior chamber paracentesis; Aqueous humor; Liquid biopsy; Low dead space; Ophthalmic needle

14
Evaluation of the Efficacy and Safety of Combination Therapy of Vamha and Myrha in the Management of PMOS: An Open-Label, Randomized, Multicentre, Comparative, Prospective Clinical Study

Patil, A.; Barathe, R.; Tate, D. M.; Kate, K.; Pande, S.; Gawande, N.; More, A.; Mahadik, S.; Berde, K.; Singhvi, R.

2026-09-02 sexual and reproductive health 10.64898/2026.08.20.26360875 medRxiv
Top 1%
0.0%
Show abstract

Introduction: Polyendocrine metabolic ovarian syndrome (PMOS), formerly known as polycystic ovary syndrome (PCOS), is a common endocrine disorder affecting women of reproductive age. Besides reproductive and metabolic disturbances, PMOS negatively impacts psychological well-being and quality of life. Despite available treatment options, there remains a need for safe and effective therapies that improve both clinical symptoms and fertility outcomes. Aim: To compare the efficacy of VAMHA and MYRHA tablet combination therapy with standard non-hormonal therapy in restoring regular menstruation. Secondary objectives included assessment of ovulation, menstrual symptoms, polycystic ovarian morphology, hormonal and metabolic parameters, anthropometric measures, and skin manifestations. Study Design: Open-label, randomized, multicentre, prospective comparative clinical study. Methods: Seventy-one women with PMOS were randomized to Group A (n=37) or Group B (n=34). Group A received VAMHA and MYRHA tablets (2 tablets each), while Group B received Metformin 500 mg plus Myoinositol 600 mg (1 tablet), twice daily for 180 days. Data were recorded in Case Report Forms. Statistical Analysis: Continuous variables were summarized using mean and standard deviation, while categorical variables were expressed as frequencies and percentages. Appropriate statistical tests, including Chi-square, were used. A p-value [&le;]0.05 was considered significant. Results: Significantly more participants in Group A achieved regular menstrual cycles than Group B (31 vs. 22; p<0.05). Ovulation occurred in 16 participants in Group A compared with 6 in Group B (p<0.05). Both groups showed significant improvement in menstrual irregularity and related symptoms. Significant reductions in Anti-Mullerian Hormone (AMH), fasting insulin, and body mass index (BMI) were observed in both groups (p<0.05). Resolution of polycystic ovarian morphology occurred in 13 participants (38.23%) in Group A and 10 (33.33%) in Group B. Both treatments were well tolerated with no major safety concerns. Conclusions: VAMHA and MYRHA combination therapy was superior to standard non-hormonal therapy in improving menstrual regularity and ovulation. It also produced favourable metabolic, hormonal, and ultrasonographic outcomes, suggesting its potential as a safe and effective option for comprehensive PMOS management and fertility enhancement.

15
Burden of fatigue in compensated chronic liver disease: findings from the multinational a:GAP Study

Choudhuri, G.; Akhundova-Unadkat, G.; Naidoo, N.; Morales-Castillo, M.; Guillaume, X.; Duijnhoven, R. G.; Safaei, A.; Swain, M. G.

2026-09-02 gastroenterology 10.64898/2026.08.28.26361618 medRxiv
Top 1%
0.0%
Show abstract

Background & Aims: Fatigue is a central symptom of chronic liver disease (CLD), substantially impacting health-related quality of life (HRQoL). This study aimed to further understand CLD symptomatology, including fatigue, and its impact on HRQoL from a patient perspective. Methods: Abbott Global Assessment of Patients unmet needs (aGAP) was a multinational, cross-sectional survey in adults with compensated CLD in China, India and Mexico, conducted between July and November 2024. Adult participants who self-reported that they had physician-diagnosed CLD and were experiencing fatigue completed a quantitative survey to assess symptom burden and included three HRQoL patient-reported outcome (PRO) questionnaires (Patient-Reported Outcomes Measurement Information System [PROMIS]-29+2, Work Productivity and Activity Impairment - Specific Health Problem version 2.0 [WPAI: SHP], Multidimensional Fatigue Inventory [MFI]). Results: Overall, 505 participants (China: 200; Mexico: 105; India: 200) completed the study. Participants reported that their CLD-related fatigue sometimes, often or always affected their self-esteem/confidence (45.1%) and ability to maintain or acquire new employment (38.6%). Most participants reported moderate (51.3%) or serious (26.9%) fatigue, with 33.5% experiencing fatigue every day or almost every day. Many participants felt their social life was negatively impacted by their fatigue (47.3%) and that there were related financial difficulties (53.9%). Use of validated PRO tools demonstrated severe fatigue (MFI: overall mean [SD] 13.9 [3.4] general fatigue and 13.4 [3.6] physical fatigue) as well as substantial levels of work and activity impairment (WPAI: SHP overall mean [SD] 53.0 [26.4]) and high levels of anxiety, pain interference, depression and sleep interference (PROMIS T-scores [&ge;]54). Conclusions: Fatigue has a substantial impact on HRQoL among adults with CLD across several countries, highlighting a global unmet need for targeted interventions to effectively identify and manage the condition.

16
An interpretable, formally verified point-of-care ultrasound risk equation for difficult videolaryngoscopy: development and internal validation

Oyarzun-Silva, R. A.; Hernandez-Hernandez, P.; Fernandez-Vaquero, M. A.; De Luis-Cabezon, N.

2026-09-02 anesthesia 10.64898/2026.08.28.26361621 medRxiv
Top 1%
0.0%
Show abstract

Background. Videolaryngoscopy still requires adjuncts or hyperangulated rescue in a clinically important minority, and bedside screening discriminates modestly. Point-of-care ultrasound (POCUS) of the anterior airway is a promising alternative, but existing prediction models are opaque or assume a pre-specified functional form. We developed and internally validated a parsimonious, fully disclosed POCUS risk equation whose form is recovered from data and whose structural properties are machine-checked by formal proof - to our knowledge the first formally verified clinical risk predictor - following TRIPOD+AI 2024. Methods. In a prospective single-centre, single-operator cohort of 259 adults undergoing elective videolaryngoscopy (no-Easy airway 68/259, 26.3%), Sequentially Thresholded Least Squares with bootstrap stability selection (B=300) screened a 71-term library of nine POCUS features and retained a seven-term logistic equation; a two-term bootstrap-stable model was pre-specified as robustness analysis. Internal validation used 5x10 repeated cross-validation plus temporal and device hold-outs, with pre-specified overfitting and optimism assessments. Five behavioural properties of the deployed equation were machine-checked in Lean 4. Results. Two interactions met the |c|/sigma_c>2 stability criterion: skin-to-epiglottis x skin-to-hyoid-bone distance and tongue volume x sagittal tongue area. The seven-term equation reached a 5x10 cross-validated C-statistic of 0.966 (optimism-corrected 0.968) and held across temporal and device hold-outs (0.94-0.97). Calibration-in-the-large matched prevalence, with cross-validated slope 0.90 attenuating to 0.625 out-of-time; standard recalibration restored 0.92 without loss of discrimination. The pre-specified two-term robustness model reproduced this performance (C-statistic 0.964-0.968; events-per-parameter 34; shrinkage 0.99), confirming the result is not an artefact of the screening stage. Net benefit over a clinical baseline was positive across 10-50% thresholds. All five Lean 4 theorems compiled without sorry. Conclusions. A sparse, formally verified POCUS equation predicts difficult videolaryngoscopy with high internally validated discrimination and quantified, modest overfitting. Because the equation was developed in a single-operator cohort and its inputs are operator-dependent, external validation requires prior harmonisation of the measurement protocol and operator credentialing.

17
The accuracy of urine-based mycobacterial antigens to detect childhood tuberculosis using an ultrasensitive immunoassay

Nkereuwem, E.; Misaghian, S.; Jaganath, D.; Calderon, R. I.; Luiz, J.; Paradkar, M.; Wambi, P.; Castro, R.; Nerurkar, R.; Wang, M.; Wohlstadter, J.; Franke, M. F.; Kampmann, B.; Kinikar, A.; Zar, H. J.; Segal, M.; Kato-Maeda, M.; Collins, J. M.; Swaney, D.; Cattamanchi, A.; Ernst, J. D.; Wobudeya, E.; Sigal, G.; The Combo Study,

2026-09-02 infectious diseases 10.64898/2026.08.28.26361530 medRxiv
Top 1%
0.0%
Show abstract

Background. Urine-based testing offers a promising non-sputum approach for diagnosing paediatric tuberculosis. However, the currently available lipoarabinomannan (LAM) assay shows limited sensitivity in children and is primarily indicated for those living with HIV. Co-detection of LAM with Mycobacterium tuberculosis (Mtb) proteins in urine could provide complementary pathogen-derived biomarkers that improve diagnostic performance. Methods. We developed an ultrasensitive multiplex electrochemiluminescence (ECL) immunoassay to measure Ag85B, CFP-10, ESAT-6, MPT32, and MPT64 in urine. We determined the analytical limits of detection and evaluated the diagnostic performance of individual proteins and LAM using urine samples from children with Confirmed, Unconfirmed, and Unlikely pulmonary tuberculosis enrolled across five high-burden countries (The Gambia, India, Peru, South Africa, and Uganda). Performance was assessed overall, by HIV and nutritional status, and across biomarker combinations. Findings. Urine samples from 630 children were analysed (median age was 4 years [IQR 2-8]; 44% female, 15% living with HIV, 19% underweight, 24% with Confirmed tuberculosis). The ECL assay achieved femtomolar limits of detection (1.5 to 4.0 fM). The sensitivity and specificity of individual Mtb proteins were 12-33% and 98-100%, respectively. Ag85B had the highest sensitivity (33%, 95% CI 26-41) for Confirmed tuberculosis and was similar to LAM. A four-antigen signature (Ag85B, MPT64, MPT32, LAM) was 50% sensitive (95% CI 42-58) and 94% specific (95% CI 90-96), and was significantly more sensitive than LAM alone, in particular among those without HIV. An additional sixteen (10%) of children with Unconfirmed TB had at least one Mtb protein or LAM detected. Interpretation. Multiple Mtb proteins are detectable in paediatric urine with high specificity, and multi-antigen signatures can augment sensitivity versus LAM alone. These findings demonstrate the potential of multi-antigen urine detection for childhood TB and define analytical targets for the development of future point-of-care diagnostics. Funding. National Institutes of Health.

18
Glaucoma and Diabetes Mellitus: A Comparative Evaluation of Comorbid Effect on Tear Quantity among Patients in Owerri, Imo State, Nigeria.

Chukwuoha, C. M.; Ovenseri-Ogbomo, G.; Azuamah, Y. C.; Odimegwu, N. E.; Obioma-Elemba, J. E.; Ugwoke, G.; Nkeremuzor, E. C.; Eronini, Y.; Ikoro, N. C.; Esenwah, E. C.

2026-09-02 ophthalmology 10.64898/2026.08.30.26361782 medRxiv
Top 1%
0.0%
Show abstract

Abstract Objective: Glaucoma is a chronic disorder that impairs ocular health and may exacerbate ocular surface disease leading to tear film instability, dry eye symptoms and decreased quality of life. This study compared changes in tear quantity among glaucoma subjects living with and without diabetes mellitus, attending an eye clinic in Nigeria. Methods: A comparative cross sectional research design was used. 157 subjects which comprised 74 glaucoma subjects living with diabetes mellitus and 83 glaucoma subjects living without diabetes mellitus participated in the study. Tear quantity assessment included the Schirmer I test and tear meniscus height (TMH) measurement. Descriptive statistics, independent samples t-test and Chi-square test were used to examine the data at 0.05 level of significance. Results: Glaucoma subjects living with diabetes mellitus showed substantially decreased tear production (11.4 +/- 6.8 mm) compared with glaucoma subjects living without diabetes mellitus (19.6 +/- 9.6 mm; p < 0.001). Tear meniscus height in glaucoma subjects living with diabetes mellitus (0.8 +/- 0.3 mm) was significantly greater than in subjects living without diabetes mellitus (0.7 +/- 0.3 mm; p = 0.034). Conclusion: Diabetes mellitus dramatically deteriorates the ocular surface function in glaucoma subjects by decreasing tear production, altering the tear meniscus height and increasing the severity of ocular surface symptoms. Routine glaucoma care, especially in patients with diabetes mellitus, should include a full ocular surface evaluation including Schirmer I test, TBUT, TMH, and OSDI assessment to allow early detection and management of ocular surface disease, better treatment adherence, and improved visual outcomes. Keywords: Glaucoma, Diabetes Mellitus, Tear production, Tear Meniscus Height, Ocular Surface Disease.

19
Non-inferior survival and enhanced longevity with initial low-dose versus full-dose enzalutamide: a single-centre real-world prostate cancer study

Gorobets, O.; Vinh-Hung, V.

2026-09-02 oncology 10.64898/2026.08.28.26361616 medRxiv
Top 1%
0.0%
Show abstract

Background: Prostate cancer enzalutamide treatment is approved at a standard dose of 160 mg daily. Concerns for real-world patients -- older and more fragile than those enrolled in clinical trials -- have prompted consideration of initiating treatment with lower doses, but the long-term efficacy of this approach remains unknown. We evaluate the long-term survival and longevity in patients treated with standard versus upfront low-dose enzalutamide. Methods: Retrospective analysis of 151 patients treated with enzalutamide (102 receiving 160 mg; 49 receiving [&le;]80 mg) between 2014--2021 at the Centre Hospitalier Universitaire de Martinique, with complete follow-up through end of life (98.7% completeness of follow-up). Primary outcomes were overall survival (OS), progression-free survival (PFS), and longevity (attained age). Results: Doses [&le;]80 mg were associated with longer median OS (36.3 vs. 20.7 months), improved restricted mean OS (difference of 0.7 years, p=0.05), and enhanced longevity (median 82.5 vs. 78.3 years, p=0.004). PSA response rate at 12 weeks was higher with lower-dose (71.4% vs. 48.8%, p=0.016). In multivariable models adjusted for prognostic factors, [&le;]40 mg compared with 160 mg was non-inferior regarding OS (HR=0.61, 95% CI 0.36--1.06), superior regarding PFS (HR=0.59, 95% CI 0.35--0.99), and superior regarding longevity (HR=0.48, 95% CI 0.28--0.84). Bone metastasis, poor performance status, PSA response, time to PSA nadir, and disease duration were independent predictors of outcomes. A post-hoc analysis revealed a strong association between dose and physician-prescribing profiles, ranging from "endorse-lowest-dose" to "never-deviate-from-full-dose". Conclusions: Lower doses of enzalutamide were non-inferior to full-dose. Dose-adapted strategies warrant further investigation.

20
Maternal cell-free RNA versus combined screening for first-trimester prediction of early-onset preeclampsia: a nested case-control study

Satorres-Perez, E.; Castillo-Marco, N.; Igual, M.; Cordero, T.; Munoz-Blat, I.; Monfort-Ortiz, R.; Marcos-Puig, B.; Simon, C.; Garrido-Gomez, T.; Perales-Marin, A.

2026-09-02 obstetrics and gynecology 10.64898/2026.08.28.26361628 medRxiv
Top 1%
0.0%
Show abstract

Background. In Europe, first-trimester combined screening with the Fetal Medicine Foundation (FMF) algorithm identifies women at increased risk of preeclampsia who may benefit from personalized aspirin prophylaxis. However, a substantial proportion of early-onset preeclampsia (EOPE) remains undetected at clinically acceptable specificity. Objective. To evaluate the first-trimester performance of MaiRa for early-onset preeclampsia (EOPE) risk stratification by benchmarking it against FMF screening in the same women, characterizing discordant patient-level classification profiles and exploring potential implementation strategies. Study Design. This secondary case-control analysis was nested within the prospective, multicentre PREMOM cohort [NCT04990141], which enrolled women with singleton pregnancies across 14 tertiary hospitals in Spain. First-trimester MaiRa and FMF risk estimates were evaluated in the same 126 pregnant women, comprising 99 uncomplicated controls and 27 EOPE cases, defined by disease onset before 34 weeks. Discrimination was compared using a stratified paired bootstrap analysis of the areas under the receiver-operating-characteristic curves. Performance was assessed at prespecified clinical thresholds, and detection rates were evaluated at fixed false-positive rates. Universal and contingent MaiRa implementation strategies were also evaluated. Results. MaiRa showed greater first-trimester discrimination for EOPE than FMF combined screening (AUC, 0.974 vs 0.900; P=.040) and consistently achieved higher detection rates across fixed false-positive rates. At false-positive rates of 5% and 10%, MaiRa detected 85.2% and 92.6% of EOPE cases, compared with 44.4% and 70.4% for FMF, respectively. Patient-level analysis demonstrated that MaiRa identified 12 of 27 EOPE cases (44.4%) classified as low risk by FMF; these pregnancies generally exhibited less abnormal conventional first-trimester profiles, including fewer maternal risk factors, lower mean arterial pressure and lower uterine artery pulsatility index, yet 8 of 12 (66.7%) subsequently developed severe EOPE. Exploratory implementation analyses showed that universal MaiRa screening achieved the highest EOPE detection, whereas a contingent strategy using FMF for triage and reflex MaiRa testing reduced molecular testing to 35.7% of pregnancies while maintaining 77.8% sensitivity and 97.0% specificity. Conclusion. MaiRa provided greater first-trimester discrimination for EOPE than conventional combined screening and detected additional pregnancies that later developed severe disease despite less abnormal conventional screening profiles. The findings suggest that maternal plasma cfRNA profiling captures biological alterations not fully reflected by combined first-trimester screening and support further prospective evaluation in an independent, unselected obstetric population. Key words: early-onset preeclampsia; first-trimester screening; cell-free RNA; liquid biopsy; Fetal Medicine Foundation algorithm; combined screening; risk stratification; aspirin prophylaxis.